NSF PAR Search | NSF Public Access Repository

Note: When clicking on a Digital Object Identifier (DOI) number, you will be taken to an external site maintained by the publisher. Some full text articles may not yet be available without a charge during the embargo (administrative interval).
What is a DOI Number?

Some links on this page may take you to non-federal websites. Their policies may differ from this site.

CaT-Bench: Benchmarking Language Model Understanding of Causal and Temporal Dependencies in Plans

Lal, Yash Kumar; Cohen, Vanya; Chambers, Nathanael; Balasubramanian, Niranjan; Mooney, Ray (November 2024, Proceedings of the 2024 Conference on Empirical Methods in Natural Language Processing)

Full Text Available
Multimodal Contextualized Semantic Parsing from Speech

https://doi.org/10.18653/v1/2024.acl-long.398

Voas, Jordan; Harwath, David; Mooney, Ray (August 2024, Association for Computational Linguistics)

Full Text Available
Analyzing the Contribution of Commonsense Knowledge Sources for Why-Question Answering

Lal, Yash Kumar; Liu, Horace; Tandon, Niket; Chambers, Nathanael; Mooney, Ray; Balasubramanian, Niranjan (May 2022, ACL 2022 Workshop on Commonsense Representation and Reasoning)

Full Text Available
Learning to Describe Solutions for Bug Reports Based on Developer Discussions

https://doi.org/10.18653/v1/2022.findings-acl.231

Panthaplackel, Sheena; Li, Junyi Jessy; Gligoric, Milos; Mooney, Ray (January 2022, Findings of the Association for Computational Linguistics: ACL 2022)

When a software bug is reported, developers engage in a discussion to collaboratively resolve it. While the solution is likely formulated within the discussion, it is often buried in a large amount of text, making it difficult to comprehend and delaying its implementation. To expedite bug resolution, we propose generating a concise natural language description of the solution by synthesizing relevant content within the discussion, which encompasses both natural language and source code. We build a corpus for this task using a novel technique for obtaining noisy supervision from repository changes linked to bug reports, with which we establish benchmarks. We also design two systems for generating a description during an ongoing discussion by classifying when sufficient context for performing the task emerges in real-time. With automated and human evaluation, we find this task to form an ideal testbed for complex reasoning in long, bimodal dialogue context.
more » « less
Full Text Available
Impact of Evaluation Methodologies on Code Summarization

https://doi.org/10.18653/v1/2022.acl-long.339

Nie, Pengyu; Zhang, Jiyang; Li, Junyi Jessy; Mooney, Ray; Gligoric, Milos (January 2022, Proceedings of the 60th Annual Meeting of the Association for Computational Linguistics)

There has been a growing interest in developing machine learning (ML) models for code summarization tasks, e.g., comment generation and method naming. Despite substantial increase in the effectiveness of ML models, the evaluation methodologies, i.e., the way people split datasets into training, validation, and test sets, were not well studied. Specifically, no prior work on code summarization considered the timestamps of code and comments during evaluation. This may lead to evaluations that are inconsistent with the intended use cases. In this paper, we introduce the time-segmented evaluation methodology, which is novel to the code summarization research community, and compare it with the mixed-project and cross-project methodologies that have been commonly used. Each methodology can be mapped to some use cases, and the time-segmented methodology should be adopted in the evaluation of ML models for code summarization. To assess the impact of methodologies, we collect a dataset of (code, comment) pairs with timestamps to train and evaluate several recent ML models for code summarization. Our experiments show that different methodologies lead to conflicting evaluation results. We invite the community to expand the set of methodologies used in evaluations.
more » « less
Full Text Available

Search for: All records